Avatar of Shangeth Rajaa

Shangeth Rajaa

Open to full-time & consulting roles in Voice AI

Senior Research Scientist working on Voice AI, Turn-Taking, Full-Duplex Spoken Dialogue Systems, and Multi-Modal Speech LLMs.

Resume
  • About
  • CV
  • Publications
  • Blog
  • Courses

#speech representation

Content tagged with "speech representation"

SpeechLLM: Multi-Modal LLM for Speech Understanding
2024-06-26
#Speech LLM #Voice AI #Speech Representation #Prosody

A small multimodal LLM that reads paralinguistic signal — emotion, prosody, speaker traits — directly from speech audio instead of through an ASR transcript, built alongside the release of SpeechLLM at Skit.ai.

Feature Disentanglement - I
2022-02-22
#Voice AI #Speech Representation #Deep Learning

How deep learning models can isolate independent factors of variation in data through VAEs and Beta-TCVAE, enabling controlled synthesis and better downstream representations.

View
Learning Speaker Representation with Semi-supervised Learning Approach for Speaker Profiling
2021-10-24 Shangeth Rajaa, Pham Van Tung, Chng Eng Siong arXiv preprint arXiv:2110.13653
#Voice AI #Speaker Profiling #Speech Representation #Semi-supervised Learning

A semi-supervised framework for speaker profiling that leverages external unlabelled corpora via supervised, unsupervised, and consistency training, achieving RMSE of 6.8 years on age estimation.

View
© 2026 Shangeth Rajaa.